- Home
- Search Results
- Page 1 of 1
Search for: All records
-
Total Resources3
- Resource Type
-
0003000000000000
- More
- Availability
-
30
- Author / Contributor
- Filter by Author / Creator
-
-
Chodroff, Eleanor (3)
-
Cotterell, Ryan (2)
-
Pimentel, Tiago (2)
-
Salesky, Elizabeth (2)
-
Anastasopoulos, Antonios (1)
-
Ben Simon, Talia (1)
-
Black, Alan W (1)
-
Chernyak, Bronya Roni (1)
-
Cole, Jennifer (1)
-
Cruz, Hilaria (1)
-
Czarnowska, Paula (1)
-
Eisner, Jason (1)
-
Hall Maudslay, Rowan (1)
-
Hulden, Mans (1)
-
Keshet, Joseph (1)
-
Kirov, Christo (1)
-
Klyachko, Elena (1)
-
Krizhanovsky, Andrew (1)
-
Krizhanovsky, Natalia (1)
-
Mielke, Sabrina J. (1)
-
- Filter by Editor
-
-
& Spizer, S. M. (0)
-
& . Spizer, S. (0)
-
& Ahn, J. (0)
-
& Bateiha, S. (0)
-
& Bosch, N. (0)
-
& Brennan K. (0)
-
& Brennan, K. (0)
-
& Chen, B. (0)
-
& Chen, Bodong (0)
-
& Drown, S. (0)
-
& Ferretti, F. (0)
-
& Higgins, A. (0)
-
& J. Peters (0)
-
& Kali, Y. (0)
-
& Ruiz-Arias, P.M. (0)
-
& S. Spitzer (0)
-
& Sahin. I. (0)
-
& Spitzer, S. (0)
-
& Spitzer, S.M. (0)
-
(submitted - in Review for IEEE ICASSP-2024) (0)
-
-
Have feedback or suggestions for a way to improve these results?
!
Note: When clicking on a Digital Object Identifier (DOI) number, you will be taken to an external site maintained by the publisher.
Some full text articles may not yet be available without a charge during the embargo (administrative interval).
What is a DOI Number?
Some links on this page may take you to non-federal websites. Their policies may differ from this site.
-
Vocal fry or creaky voice refers to a voice quality characterized by irregular glottal opening and low pitch. It occurs in diverse languages and is prevalent in American English, where it is used not only to mark phrase finality, but also sociolinguistic factors and affect. Due to its irregular periodicity, creaky voice challenges automatic speech processing and recognition systems, particularly for languages where creak is frequently used. This paper proposes a deep learning model to detect creaky voice in fluent speech. The model is composed of an encoder and a classifier trained together. The encoder takes the raw waveform and learns a representation using a convolutional neural network. The classifier is implemented as a multi-headed fully-connected network trained to detect creaky voice, voicing, and pitch, where the last two are used to refine creak prediction. The model is trained and tested on speech of American English speakers, annotated for creak by trained phoneticians. We evaluated the performance of our system using two encoders: one is tailored for the task, and the other is based on a state-of-the-art unsupervised representation. Results suggest our best-performing system has improved recall and F1 scores compared to previous methods on unseen data.more » « less
-
Salesky, Elizabeth; Chodroff, Eleanor; Pimentel, Tiago; Wiesner, Matthew; Cotterell, Ryan; Black, Alan W; Eisner, Jason (, Proceedings of the 58th Annual Meeting of the Association for Computational Linguistics (ACL))
-
Vylomova, Ekaterina; White, Jennifer; Salesky, Elizabeth; Mielke, Sabrina J.; Wu, Shijie; Ponti, Edoardo Maria; Hall Maudslay, Rowan; Zmigrod, Ran; Valvoda, Josef; Toldova, Svetlana; et al (, Proceedings of the 17th SIGMORPHON Workshop on Computational Research in Phonetics, Phonology, and Morphology)
An official website of the United States government
